Papers with recognizing disguised offensive language
Fortifying Toxic Speech Detectors Against Veiled Toxicity (2020.emnlp-main)
Copied to clipboard
| Challenge: | Modern toxic speech detectors are incompetent in recognizing disguised offensive language, such as adversarial attacks that deliberately avoid known toxic lexicons. |
| Approach: | They propose a framework that fortifies existing toxic speech detectors without a large labeled corpus of veiled toxicity. |
| Outcome: | The proposed framework is aimed at fortifying existing toxic speech detectors without a large labeled corpus of disguised offensive language. |